Goto

Collaborating Authors

 flag hateful content


Want AI that flags hateful content? Build it.

MIT Technology Review

The challenge asks for two different models. The first, a task for those with intermediate skills, is one that identifies hateful images; the second, considered an advanced challenge, is a model that attempts to fool the first one. "That actually mimics how it works in the real world," says Chowdhury. "The do-gooders make one approach, and then the bad guys make an approach." The goal is to engage machine-learning researchers on the topic of mitigating extremism, which may lead to the creation of new models that can effectively screen for hateful images.